Homogeneity and change-point detection tests for multivariate data using rank statistics

نویسندگان

  • Alexandre Lung-Yut-Fong
  • Céline Lévy-Leduc
  • Olivier Cappé
چکیده

Detecting and locating changes in highly multivariate data is a major concern in several current statistical applications. In this context, the first contribution of the paper is a novel non-parametric two-sample homogeneity test for multivariate data based on the well-known Wilcoxon rank statistic. The proposed two-sample homogeneity test statistic can be extended to deal with ordinal or censored data as well as to test for the homogeneity of more than two samples. We also provide a detailed analysis of the power of the proposed test statistic (in the two sample case) against asymptotic local shift alternatives. The second contribution of the paper concerns the use of the proposed test statistic to perform retrospective change-point detection. It is first shown that the approach is computationally feasible even when looking for a large number of change-points thanks to the use of dynamic programming. Computable asymptotic p-values for the test are available in the case where a single potential change-point is to be detected. The proposed approach is particularly recommendable in situations where the correlations between the coordinates of the data are moderate, the marginal distributions are not well modelled by usual parametric assumptions (e.g., in the presence of outliers) and when faced with highly variable change patterns, for instance, if the potential changes only affect subsets of the coordinates of the data.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Nonparametric tests for change-point detection à la Gombay and Horváth

The nonparametric test for change-point detection proposed by Gombay and Horváth is revisited and extended in the broader setting of empirical process theory. The resulting testing procedure for potentially multivariate observations is based on a sequential generalization of the functional multiplier central limit theorem and on modifications of Gombay and Horváth’s seminal approach that appear...

متن کامل

Detecting and estimating changes in dependent functional data

Change point detection in sequences of functional data is examined where the functional observations are dependent. Of particular interest is the case where the change point is an epidemic change (a change occurs and then the observations return to baseline at a later time). The theoretical properties for various tests for at most one change and epidemic changes are derived with a special focus...

متن کامل

Identification of outliers types in multivariate time series using genetic algorithm

Multivariate time series data, often, modeled using vector autoregressive moving average (VARMA) model. But presence of outliers can violates the stationary assumption and may lead to wrong modeling, biased estimation of parameters and inaccurate prediction. Thus, detection of these points and how to deal properly with them, especially in relation to modeling and parameter estimation of VARMA m...

متن کامل

On the Canonical-Based Goodness-of-fit Tests for Multivariate Skew-Normality

It is well-known that the skew-normal distribution can provide an alternative model to the normal distribution for analyzing asymmetric data. The aim of this paper is to propose two goodness-of-fit tests for assessing whether a sample comes from a multivariate skew-normal (MSN) distribution. We address the problem of multivariate skew-normality goodness-of-fit based on the empirical Laplace tra...

متن کامل

Change Points via Probabilistically Pruned Objectives

The concept of homogeneity plays a critical role in statistics, both in its applications as well as its theory. Change point analysis is a statistical tool that aims to attain homogeneity within time series data. This is accomplished through partitioning the time series into a number of contiguous homogeneous segments. The applications of such techniques range from identifying chromosome altera...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 2012